Papers by Ali Sartaz Khan
AudioJudge: Understanding What Works in Large Audio Model Based Speech Evaluation (2026.eacl-long)
Copied to clipboard
Potsawee Manakul, Woody Haosheng Gan, Michael J Ryan, Ali Sartaz Khan, Warit Sirichotedumrong, Kunat Pipatanakul, William Barr Held, Diyi Yang
| Challenge: | Current speech evaluation systems rely on specialized systems for individual audio characteristics and poor correlation between automatic methods and human preferences. |
| Approach: | They propose a unified evaluation framework for Large Audio Models as a Judge, AudioJudge . they propose specialized judges that can be prompted to perform audio characteristic detection tasks . |
| Outcome: | The proposed method improves performance across audio characteristic detection and human preference simulation tasks. |